Papers with Automatic text generation

2 papers
Can Synthetic Text Help Clinical Named Entity Recognition? A Study of Electronic Health Records in French (2023.eacl-main)

Copied to clipboard

Challenge: In sensitive domains, the sharing of corpora is restricted due to confidentiality, copyrights or trade secrets.
Approach: They use auto-regressive neural models to generate a clinical case corpus annotated with clinical entities and evaluate it for a named entity recognition task.
Outcome: The proposed model can produce clinical case corpus annotated with clinical entities while maintaining confidentiality.
A Benchmark Corpus for the Detection of Automatically Generated Text in Academic Publications (2022.lrec-1)

Copied to clipboard

Challenge: Automated text generation has achieved performance levels that make the generated text almost indistinguishable from those written by humans.
Approach: They propose to use a completely synthetic dataset and a partial text substitution dataset to evaluate the quality of the generated research content.
Outcome: The proposed datasets compare the generated texts to aligned original texts using fluency metrics such as BLEU and ROUGE.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations